AITopics | Eastern Region

Collaborating Authors

Eastern Region

A Cookbook for Community-driven Data Collection of Impaired Speech in LowResource Languages

Salihs, Sumaya Ahmed, Wiafe, Isaac, Abdulai, Jamal-Deen, Atsakpo, Elikem Doe, Ayoka, Gifty, Cave, Richard, Ekpezu, Akon Obu, Holloway, Catherine, Tomanek, Katrin, Winful, Fiifi Baffoe Payin

arXiv.org Artificial IntelligenceJul-4-2025

This study presents an approach for collecting speech samples to build Automatic Speech Recognition (ASR) models for impaired speech, particularly, low-resource languages. It aims to democratize ASR technology and data collection by developing a "cookbook" of best practices and training for community-driven data collection and ASR model building. As a proof-of-concept, this study curated the first open-source dataset of impaired speech in Akan: a widely spoken indigenous language in Ghana. The study involved participants from diverse backgrounds with speech impairments. The resulting dataset, along with the cookbook and open-source tools, are publicly available to enable researchers and practitioners to create inclusive ASR technologies tailored to the unique needs of speech impaired individuals. In addition, this study presents the initial results of fine-tuning open-source ASR models to better recognize impaired speech in Akan.

artificial intelligence, participant, speech recognition, (14 more...)

arXiv.org Artificial Intelligence

2507.02428

Country:

North America > United States (0.14)
Europe > United Kingdom > England > Greater London > London (0.14)
Africa > Sub-Saharan Africa (0.05)
(4 more...)

Genre: Research Report > New Finding (1.00)

Industry: Health & Medicine > Therapeutic Area > Neurology (1.00)

Technology: Information Technology > Artificial Intelligence > Speech > Speech Recognition (1.00)

Add feedback

Vernacular? I Barely Know Her: Challenges with Style Control and Stereotyping

Aich, Ankit, Liu, Tingting, Giorgi, Salvatore, Isman, Kelsey, Ungar, Lyle, Curtis, Brenda

arXiv.org Artificial IntelligenceJun-18-2024

Large Language Models (LLMs) are increasingly being used in educational and learning applications. Research has demonstrated that controlling for style, to fit the needs of the learner, fosters increased understanding, promotes inclusion, and helps with knowledge distillation. To understand the capabilities and limitations of contemporary LLMs in style control, we evaluated five state-of-the-art models: GPT-3.5, GPT-4, GPT-4o, Llama-3, and Mistral-instruct-7B across two style control tasks. We observed significant inconsistencies in the first task, with model performances averaging between 5th and 8th grade reading levels for tasks intended for first-graders, and standard deviations up to 27.6. For our second task, we observed a statistically significant improvement in performance from 0.02 to 0.26. However, we find that even without stereotypes in reference texts, LLMs Figure 1: Overall view of this paper. We find that while often generated culturally insensitive content in-context learning can control for reading level and during their tasks. We provide a thorough analysis simplicity, it cannot do the same for vernacular English.

instruction, reading level, stereotype, (17 more...)

arXiv.org Artificial Intelligence

2406.12679

Country:

North America > United States > Texas > Travis County > Austin (0.04)
North America > United States > Pennsylvania (0.04)
North America > United States > Nebraska > Lancaster County > Lincoln (0.04)
(5 more...)

Genre: Research Report > New Finding (1.00)

Industry:

Law (1.00)
Health & Medicine > Therapeutic Area > Psychiatry/Psychology (0.69)
Education > Educational Setting > K-12 Education > Primary School (0.35)

Technology:

Information Technology > Artificial Intelligence > Natural Language > Large Language Model (1.00)
Information Technology > Artificial Intelligence > Natural Language > Chatbot (1.00)
Information Technology > Artificial Intelligence > Machine Learning > Neural Networks > Deep Learning (1.00)

Add feedback